Papers by Atul Kumar Singh
Samayik: A Benchmark and Dataset for English-Sanskrit Translation (2024.lrec-main)
Copied to clipboard
Ayush Maheshwari, Ashim Gupta, Amrith Krishna, Atul Kumar Singh, Ganesh Ramakrishnan, Anil Kumar Gourishetty, Jitin Singla
| Challenge: | Existing Sanskrit corpora focus on poetry and offer limited coverage of contemporary written materials. |
| Approach: | They release a dataset of 53,000 parallel English-Sanskrit sentences . they use spoken content that covers contemporary world affairs and interpretations . |
| Outcome: | a new dataset of 53,000 parallel English-Sanskrit sentences is released . the dataset outperforms existing models trained on older classical-era poetry datasets . |
LexGen: Domain-aware Multilingual Lexicon Generation (2025.acl-long)
Copied to clipboard
Ayush Maheshwari, Atul Kumar Singh, N J Karthika, Krishnakant Bhatt, Preethi Jyothi, Ganesh Ramakrishnan
| Challenge: | Lexicon generation is a key task in specialized domains due to infrequent usage of terms . a new model is proposed to generate dictionary words for 6 Indian languages . |
| Approach: | They propose a model to generate dictionary words for 6 Indian languages in the multi-domain setting. |
| Outcome: | The proposed model generalizes to unseen domains and unsealed languages. |